Retrieval-Augmented Claude Opus 4.7 and GPT-5.5 Surpass Human Performance on the Nuclear Cardiology Board Preparation Exam (and Claude Drafts a Paper About it)
下一代大型语言模型,特别是配备利用领域特异性核心脏病学资源进行检索增强生成的 Claude Opus 4.7 和 GPT-5.5,在 ASNC 董事会备考考试中实现了约 86% 的平均准确率,超越了预估的及格阈值以及人类受训研究员的平均表现。